AITopics | thompson sampling and approximate inference

Collaborating Authors

thompson sampling and approximate inference

Information about AI from the News, Publications, and Conferences

Automatic Classification – Tagging and Summarization – Customizable Filtering and Analysis

If you are looking for an answer to the question What is Artificial Intelligence? and you only have a minute, then here's the definition the Association for the Advancement of Artificial Intelligence offers on its home page: "the scientific understanding of the mechanisms underlying thought and intelligent behavior and their embodiment in machines."

However, if you are fortunate enough to have more than a minute, then please get ready to embark upon an exciting journey exploring AI (but beware, it could last a lifetime) …

Thompson Sampling and Approximate Inference

Neural Information Processing SystemsDec-26-2025, 03:31:56 GMT

electronic proceedings, name change, thompson sampling and approximate inference

Neural Information Processing Systems

Technology: Information Technology > Artificial Intelligence (0.50)

Add feedback

Reviews: Thompson Sampling and Approximate Inference

Neural Information Processing SystemsFeb-5-2025, 08:24:17 GMT

Thompson Sampling with Approximate Inference This paper investigates the performance of Thompson sampling, when the sampled distribution does not match the "problem" distribution exactly. The authors clearly explain some settings where mismatched sampling distributions can cause linear regret. The authors support their analysis with some expository "toy" experiments. There are several things to like about this paper: - This paper is one of the first to provide a clear analysis of Thompson sampling in the regime of imperfect inference. It seems like another solution that would "intuitively work" is to artificially expand the "prior" of the Thompson sampling procedure, but in a way that would concentrate away with data.

author support, review, thompson sampling and approximate inference, (2 more...)

Neural Information Processing Systems

Genre: Research Report (0.79)

Technology: Information Technology > Artificial Intelligence > Representation & Reasoning > Uncertainty (0.65)

Add feedback

Reviews: Thompson Sampling and Approximate Inference

Neural Information Processing SystemsFeb-5-2025, 08:24:06 GMT

This submission was diligently evaluated, and the main problem that was identified was clarity. An additional expert reviewer was asked to checked the proofs. I am suggesting acceptance, given that the authors put a serious effort into rewriting. The reviewers did a great job in identifying where the improvements are needed, please do take advantage of that. Furthermore, for the final version, please take a look at this paper: https://link.springer.com/chapter/10.1007%2F978-3-319-46379-7_22 .

review, reviewer, thompson sampling and approximate inference

Neural Information Processing Systems

Technology: Information Technology > Artificial Intelligence > Representation & Reasoning > Uncertainty (0.40)

Add feedback

Thompson Sampling and Approximate Inference

Neural Information Processing SystemsOct-11-2024, 05:55:18 GMT

We study the effects of approximate inference on the performance of Thompson sampling in the k -armed bandit problems. Thompson sampling is a successful algorithm for online decision-making but requires posterior inference, which often must be approximated in practice. We show that even small constant inference error (in \alpha -divergence) can lead to poor performance (linear regret) due to under-exploration (for \alpha 1) or over-exploration (for \alpha 0) by the approximation. While for \alpha 0 this is unavoidable, for \alpha \leq 0 the regret can be improved by adding a small amount of forced exploration even when the inference error is a large constant.

alpha 0, inference error, thompson sampling and approximate inference

Neural Information Processing Systems

Technology: Information Technology > Artificial Intelligence > Representation & Reasoning > Uncertainty (0.68)

Add feedback

Thompson Sampling and Approximate Inference

Phan, My, Yadkori, Yasin Abbasi, Domke, Justin

Neural Information Processing SystemsMar-19-2020, 00:04:28 GMT

We study the effects of approximate inference on the performance of Thompson sampling in the $k$-armed bandit problems. Thompson sampling is a successful algorithm for online decision-making but requires posterior inference, which often must be approximated in practice. We show that even small constant inference error (in $\alpha$-divergence) can lead to poor performance (linear regret) due to under-exploration (for $\alpha 1$) or over-exploration (for $\alpha 0$) by the approximation. While for $\alpha 0$ this is unavoidable, for $\alpha \leq 0$ the regret can be improved by adding a small amount of forced exploration even when the inference error is a large constant. Papers published at the Neural Information Processing Systems Conference.

alpha 0, inference error, thompson sampling and approximate inference

Neural Information Processing Systems

Technology: Information Technology > Artificial Intelligence > Representation & Reasoning > Uncertainty (0.68)

Add feedback